Back

Statistical Methods in Medical Research

SAGE Publications

Preprints posted in the last 7 days, ranked by how well they match Statistical Methods in Medical Research's content profile, based on 11 papers previously published here. The average preprint has a 0.01% match score for this journal, so anything above that is already an above-average fit.

1
Addressing Measurement Error of Machine-Learned Physical Activity in Nonlinear Dose-Response Survival Analysis: Development and Evaluation of Accelerated Failure Time, Spline, and Simulation-Extrapolation Method

Mamiya, H.; Zhang, Q.; Zhang, X.; Yan, Y.; Sharma, A.

2026-08-31 epidemiology 10.64898/2026.08.25.26361155 medRxiv
Top 0.3%
0.4%
Show abstract

Wearable (accelerometer) data and machine-learning allow objective assessment of the amount of daily physical activity. However, wearable-derived human activity is subject to measurement error. No studies have corrected the dose-response association between physical activity and survival time to chronic diseases, including cardiovascular disease (CVD). The objective is to estimate the measurement error-corrected association between CVD events and multiple measures of daily duration of light and total physical activity, derived from machine-learning and conventional accelerometer-processing methods. Our method combined an accelerated failure time model, spline, and simulation-extrapolation (SIMEX). The method recovered the true dose-response non-linear association in simulated data, while the naive model failed to capture it due to substantial attenuation. Application to the UK Biobank accelerometer cohort also showed an increased protective association of total physical activity after SIMEX correction (Time Ratio [TR] = 1.56, 95% CI: 1.28-1.82 vs. TR = 1.38, 95% CI: 1.24-1.54 for SIMEX-corrected vs. uncorrected dose-response association between the 95th and 5th percentiles of total activity), with a similar increase for light physical activity. Sensitivity analysis indicates that the female population experiences a substantially larger protective association after SIMEX correction than males. Dose-response survival analysis is a widely used analytical method in physical activity epidemiology and benefits from measurement error correction.

2
Bayesian Borrowing of External Information in Clinical Trials: A Comparison of MAP, RMAP, and SAM Priors

Choi, L.; McNeer, E.; Beck, C. A.; Neul, J. L.

2026-08-31 pharmacology and therapeutics 10.64898/2026.08.26.26360843 medRxiv
Top 0.4%
0.4%
Show abstract

Bayesian borrowing of external information can improve trial efficiency, particularly in pediatric and rare disease settings where patient populations are limited, but may introduce bias and inflate the Type~I error rate when the trial differs from external studies. Recent U.S. Food and Drug Administration (FDA) draft Bayesian guidance emphasizes careful evaluation of external information, prior specification, and assessment of operating characteristics. This paper compares three meta-analytic-predictive (MAP)-based methods for Bayesian borrowing: the MAP prior, robust MAP (RMAP) prior, and self-adapting mixture (SAM) prior. An adaptive platform trial design in Rett syndrome is used as a case study. Simulation studies evaluate frequentist operating characteristics under varying prior--data conflict, between-study heterogeneity, treatment effects, and clinically significant differences (CSDs) for the SAM prior. The MAP prior achieved the greatest efficiency when external and current data were compatible but exhibited the largest bias under substantial prior--data conflict. The RMAP priors improved robustness through fixed robust-component weights, whereas the SAM prior adaptively adjusted borrowing and was less sensitive to prior--data conflict while retaining efficiency gains when the data were compatible. Although the CSD influenced the degree of adaptive borrowing, as reflected by effective sample size, it had only a modest impact on frequentist operating characteristics. Sensitivity analyses using a skeptical robust component yielded similar qualitative conclusions, while accentuating the differences between the MAP and RMAP priors. These findings provide guidance for evaluating and selecting MAP-based borrowing strategies before trial implementation, particularly in rare disease settings, consistent with current FDA recommendations.

3
Genetic Architecture and Sample Size Impact Relative Performance of Nonlinear Machine Learning and Standard Polygenic Risk Scores

Zhu, J.; Baousi, A.; Morris, A. P.; Guo, H.

2026-09-03 genetic and genomic medicine 10.64898/2026.08.29.26361109 medRxiv
Top 0.6%
0.3%
Show abstract

Standard polygenic risk scores (PRSs) are constructed based on additive genome-wide association study (GWAS) summary statistics. Nonlinear machine learning methods have been increasingly applied to construct PRSs directly from individual-level data, with the aim of improving predictive performance over standard PRSs through their ability to model non-additive genetic effects. However, their superiority across studies has been inconsistent, and the conditions under which they provide meaningful improvements remain unclear. We combined theoretical analysis, simulations and a real-world application to investigate when two widely used nonlinear machine learning methods, random forest and XGBoost, outperform standard PRSs. Theoretical analysis showed that standard PRSs can implicitly capture part of the genetic variance attributable to nonadditive genetic effects through their contributions to marginal SNP effects, thereby losing less information than commonly assumed. Although nonlinear models have a higher theoretical potential, their greater flexibility incurs a bias-variance trade-off that can limit predictive gains at finite sample sizes. Simulations showed that XGBoost outperformed the standard PRS only when the genetic architecture involves a sufficiently large proportion of interaction genetic variance concentrated across relatively few interaction effects and large training samples were available. Random forest consistently underperformed the standard PRS. In an application to ischemic heart disease prediction using UK Biobank data, XGBoost showed no meaningful improvement in predictive performance over the standard PRS, whereas random forest again performed worse. Together, these findings suggest that nonlinear machine learning do not uniformly outperform standard PRSs; rather, their relative performance depends jointly on genetic architecture and training sample size. Our study helps to reconcile the inconsistent results reported across previous studies and provides a framework for identifying settings in which more complex PRS models are likely to be beneficial.

4
Primary Care Quality and Inappropriate Community Antibiotic Use: A Double Machine Learning Instrumental Variable Approach

Chen, Y.; Yi, H.; Rao, S.; Weber, A.; Hassmiller-Lich, K.; Sylvia, S.

2026-08-31 health economics 10.64898/2026.08.26.26361459 medRxiv
Top 0.6%
0.2%
Show abstract

Inappropriate antibiotic use presents a major global health challenge, particularly in low-resource settings where access to quality care is limited but antibiotics remain relatively unrestricted. This study estimates the causal effect of frontline primary care quality on inappropriate community antibiotic use, combining detailed community-based data from approximately 100 rural villages in rural China with an instrumental variable (IV) approach embedded within a double/debiased machine learning (DML) framework. We linked objective measures of village doctor clinical practice quality, measured through unannounced standardized patient visits, to household-level antibiotic use data collected from the same villages. To identify the causal effect, we constructed multiple candidate instruments from extensive provider characteristics and used an ensemble of machine learning algorithms within a flexible DML-IV framework to approximate an optimal instrument, addressing a many-weak-instruments problem. We found that improving village provider clinical practice quality reduced both antibiotic receipt during healthcare encounters for common diseases and household antibiotic storage for future self-medication. Our findings suggest that strengthening frontline primary care quality can meaningfully reduce inappropriate community antibiotic use without restricting access to essential treatment. More broadly, this study illustrates how causal machine learning can strengthen conventional causal estimation in complex observational settings in global health economics research.

5
LLM-assisted evidence audit of late-stage cancer incidence as a screening trial endpoint

Li, S.; Zhang, W.; Xing, X.; Shen, Z.; Wang, Y.; Chen, Z.; Neto, O.; Yu, Y.; Wu, C.; Lin, L.

2026-08-31 oncology 10.64898/2026.08.29.26361733 medRxiv
Top 1%
0.1%
Show abstract

Background Late-stage cancer incidence is being considered as an earlier endpoint in cancer-screening trials, but its trial-level association with cancer-specific mortality may depend on evidence selection and endpoint harmonization. We evaluated the robustness of this association to source-verified additions. Methods We reconstructed the PubMed corpus underlying a 41-comparison review. Gemini 3.1 Pro Preview was used only to prioritize reports for blinded human reassessment. Reviewers determined eligibility, linked reports from the same trial, harmonized endpoints, and verified comparison-level data. We recalculated unweighted Pearson correlations overall and by cancer type after adding earliest-compatible trial comparisons. Results Among 1209 candidate records, 996 PDFs were assessed. Thirty-three reports absent from the source review were prioritized; 26 were eligible, representing 18 trials, and 8 provided compatible comparisons. Adding these comparisons increased the dataset from 41 to 49 and attenuated the overall correlation from 0.73 (95% confidence interval [CI] = 0.55 to 0.85) to 0.59 (95% CI = 0.37 to 0.75). Updated correlations were 0.49 (95% CI = -0.26 to 0.87) for breast, -0.23 (95% CI = -0.71 to 0.40) for colorectal, and 0.83 (95% CI = 0.54 to 0.95) for lung cancer. One sparse-event comparison influenced the colorectal estimate. Conclusions The overall association was sensitive to evidence composition, and cancer-specific stability varied. Late-stage incidence should be evaluated by cancer type and with prespecified sensitivity analyses for evidence selection and endpoint definitions. Model-assisted prioritization cannot replace human eligibility review, trial reconciliation, and source verification.

6
Projected Population-Level Impact of Digital Return of Results for Cardiovascular-Kidney-Metabolic Screening at US Blood Donation Centers: A Monte Carlo Simulation Study

Qian, Z.; Khera, A.; Makhnoon, S.; Chapman, B. E.; Bryant, B.; Sayers, M.; Compton, F.; Eason, S.; Xing, C.; Ahmad, Z.

2026-09-03 public and global health 10.64898/2026.09.01.26360806 medRxiv
Top 1%
0.1%
Show abstract

Background. Cardiovascular-kidney-metabolic (CKM) syndrome affects nearly 90% of US adults, yet most individuals at early, modifiable stages remain unidentified outside clinical care. Blood donation centers offer a scalable, non-clinical venue for CKM screening, but the potential benefit of screening in this context remains unclear. We projected the population-level impact of effective digital return of results (ROR) to inform the design of a pragmatic trial. Methods. We developed a Monte Carlo simulation (100,000 iterations) of the incident major adverse cardiovascular events (MACE), end-stage renal disease (ESRD), and type 2 diabetes (T2DM) preventable by ROR-prompted, guideline-concordant follow-up among donors in CKM Stages 1-2. The estimand counts only events averted by donors who act because of ROR; the intervention effect was modeled directly on strictly positive support, and action was translated into prevented events through a hazard-based cumulative-incidence difference that counts each donor at most once. We evaluated 18 design cells (donor volumes 300,000, 1 million, and 8 million/year; 5- and 10-year horizons; action-rate gains of +10, +20, and +30 percentage points [pp]) and, in a complementary two-arm simulation, the assurance (expected power) of detecting the effect in a single deployment. Results. Under the primary +20 pp scenario, ROR at a single large blood center (300,000 donors/year) is projected to prevent a median of 2,201 events (95% uncertainty interval [UI], 1,099-4,364) over 10 years, scaling to 58,526 (29,154-116,769) at the national donor pool. All 18 design cells had strictly positive 95% lower bounds. The number needed to screen was 136 and the screening cost $2,045 per event prevented (at $15/donor), both invariant to donor volume. Impact scaled linearly with volume and effect size but sub-linearly with the horizon. Detection of the effect was effectively certain at gains of +20 pp or larger (assurance [≥]99.6% in every cell and >99.9% in all but the smallest 5-year cell). Conclusions. Even under the conservative scenario, digital CKM ROR at blood donation centers is projected to prevent hundreds to tens of thousands of incident cardiometabolic events at a screening cost per event well within accepted prevention benchmarks, providing prospective, quantitative justification for a pragmatic, randomized evaluation of digital ROR in non-clinical screening settings.

7
Non-inferior survival and enhanced longevity with initial low-dose versus full-dose enzalutamide: a single-centre real-world prostate cancer study

Gorobets, O.; Vinh-Hung, V.

2026-09-02 oncology 10.64898/2026.08.28.26361616 medRxiv
Top 1%
0.1%
Show abstract

Background: Prostate cancer enzalutamide treatment is approved at a standard dose of 160 mg daily. Concerns for real-world patients -- older and more fragile than those enrolled in clinical trials -- have prompted consideration of initiating treatment with lower doses, but the long-term efficacy of this approach remains unknown. We evaluate the long-term survival and longevity in patients treated with standard versus upfront low-dose enzalutamide. Methods: Retrospective analysis of 151 patients treated with enzalutamide (102 receiving 160 mg; 49 receiving [≤]80 mg) between 2014--2021 at the Centre Hospitalier Universitaire de Martinique, with complete follow-up through end of life (98.7% completeness of follow-up). Primary outcomes were overall survival (OS), progression-free survival (PFS), and longevity (attained age). Results: Doses [≤]80 mg were associated with longer median OS (36.3 vs. 20.7 months), improved restricted mean OS (difference of 0.7 years, p=0.05), and enhanced longevity (median 82.5 vs. 78.3 years, p=0.004). PSA response rate at 12 weeks was higher with lower-dose (71.4% vs. 48.8%, p=0.016). In multivariable models adjusted for prognostic factors, [≤]40 mg compared with 160 mg was non-inferior regarding OS (HR=0.61, 95% CI 0.36--1.06), superior regarding PFS (HR=0.59, 95% CI 0.35--0.99), and superior regarding longevity (HR=0.48, 95% CI 0.28--0.84). Bone metastasis, poor performance status, PSA response, time to PSA nadir, and disease duration were independent predictors of outcomes. A post-hoc analysis revealed a strong association between dose and physician-prescribing profiles, ranging from "endorse-lowest-dose" to "never-deviate-from-full-dose". Conclusions: Lower doses of enzalutamide were non-inferior to full-dose. Dose-adapted strategies warrant further investigation.

8
Can Dental AI Really Beat Dentists? DentalPair-Cert for Rigorous AI-Dentist Inference

Alve, S. R.; Rahman, S.; Meem, S. M. A. C.

2026-09-02 dentistry and oral medicine 10.64898/2026.09.01.26361874 medRxiv
Top 1%
0.1%
Show abstract

A dental AI system and a dentist reading the same radiographs form a paired comparison. Published comparative studies often report the two arms separately against a reference standard, leaving the joint pattern of correctness between them unavailable for secondary paired inference. We show what that omission costs. The accuracy difference remains exactly identified; its sampling variance does not, so the report contains the estimate and not its uncertainty. On a study of 282 units, two published accuracies are consistent with 38 distinct joint tables whose confidence intervals differ in width by a factor of 2.5. The consequence is a three-zone decision map rather than a single threshold: differences at or below 1.06 points are non-significant under every compatible table, differences at or above 6.03 points are significant under every compatible table, and in between the published numbers cannot decide. We then show the omission is repairable at negligible cost. One additional integer, the number of units both arms classify correctly, identifies the joint table exactly and restores standard paired inference. For a panel of readers the pairwise dependences must arise from one joint distribution, a constraint that binds once three readers are present; publishing each reader's joint-correct count against a single reference reader cannot widen and may tighten every pairwise bound, and in a 7-arm experiment reduced them by a median of 37% even for pairs excluding that reference. Where the integer was never published we give DentalPair-Cert, an interval with finite-sample coverage uniformly over every admissible within-unit AI-dentist dependence under the independent-sampling-unit model, certified in both the nuisance maximization and the inversion. Across 4,200,000 simulated comparisons an independence analysis falls to 74.5% coverage with 12.2% type-I error; in a purposive sample of 9 recent comparative studies, 1 reported a paired test on discordant units.

9
Policy implementation gap: a qualitative evaluation of the implementation of Ghana's national cancer control strategy

Jawhara, B.; Baatiema, L.

2026-08-31 health policy 10.64898/2026.08.28.26361597 medRxiv
Top 2%
0.0%
Show abstract

Background: Cancer is a growing public health challenge in Ghana, with 27,385 new cases and 17,944 deaths recorded in 2022. Ghana developed a National Cancer Control Strategy (NCCS) in 2011 to guide prevention, early detection, treatment, and palliative care. The strategy expired in 2016 and has not been formally evaluated or renewed, leaving cancer control efforts without a guiding policy framework for nearly a decade. This study examined how the strategy was implemented, what barriers were encountered and what stakeholders recommend for a strengthened national cancer response. Methods: We conducted a qualitative descriptive study using semi-structured key informant interviews. Fifteen participants were recruited through purposive sampling, supplemented by snowball referrals, representing three groups: Ministry of Health policymakers, frontline healthcare providers and representatives of cancer-focused non-governmental organisations. Data were collected between June and September 2025 and analysed using Braun and Clarke's six-phase thematic analysis framework, guided deductively by the WHO Health Systems Building Blocks framework Results: Three themes emerged: NCCS interventions and systems implemented, capturing progress in cancer awareness, HPV vaccination and pilot screening programmes alongside persistent geographic and financial inequities in access; barriers to implementation, including inadequate financing, infrastructure and workforce shortages, the absence of a national cancer registry and governance failures, among them the finding that no frontline healthcare provider interviewed had any awareness of the NCCS; and recommended implementation strategies, including co-production of a renewed strategy, establishment of a dedicated National Cancer Control Programme, expanded health insurance coverage and decentralisation of oncology services. Conclusion: The NCCS was not operationally embedded in the health system. The evidence points to failures in policy dissemination as a constraint that precedes resource constraints. Addressing Ghana's rising cancer burden requires renewed political commitment, co-produced governance structures and accountability mechanisms. These findings have relevance for other low- and middle-income country settings facing similar challenges.

10
Impact of Detection-Isolation-Leakage on the 2026 DRC Bundibugyo Ebolavirus Outbreak

Oraby, T.; Falay, D.; Ndeffo-Mbah, M. L.

2026-08-31 public and global health 10.64898/2026.08.25.26361360 medRxiv
Top 2%
0.0%
Show abstract

The 17th Ebola outbreak in the Democratic Republic of the Congo, announced on 15 May 2026, was attributed to Bundibugyo ebolavirus (BDBV). Although case isolation is the main control strategy, its effectiveness is compromised when patients escape isolation facilities before recovery. Between 14 May and 17 June 2026, 175 individuals reportedly left isolation facilities without formal discharge across Ituri Province. We assessed how this "isolation leakage" affects community transmission. We refined the SEIHFR framework to distinguish undetected community infections, detected but not-yet-isolated cases, isolated individuals, leakage, funeral-associated transmission, and removals. Using Bayesian inference, we fitted the model to daily Ituri surveillance data, escapee counts, and isolation census records. We estimated the leakage rate, reporting and detection probabilities, and the transmission rate, while fixing other parameters based on the BDBV literature. The model reproduced confirmed cases, deaths, discharges, and escapees. We estimated R_0=3.67 (95% HDI: 2.0-5.7), a leakage rate of {rho} {approx} 0.034 day^-1 (0.022-0.051), and high contact-tracing-driven detection (p_d {approx} 0.91-0.99). Leakage increased the detection-dependent reproduction number [R](p_d) from approximately 3.2 to above 5. Eliminating leakage reduced cumulative infections by about one-third, from 1,120 to 764, while the minimum detection level required for control increased from p_d [≥] 0.73 without leakage to p_d [≥] 0.87 at the fitted leakage rate. Shortening time to isolation prevented the most infections (73.4%; 59-84), followed by reducing leakage (29.7%; 14-52) and re-isolating escapees (12.6%; 6-24). Delaying leakage reduction until week 4 reduced its benefit from about 27% to below 2%. Isolation leakage represents a major transmission pathway that has until now gone largely unmeasured. While rapid initiation of isolation is highly beneficial, it cannot compensate for permeable isolation; therefore, early, community-driven efforts to control leakage, embedded within a multilayered response, are critical.

11
Community health system vital signs and preventable neonatal mortality in Mashonaland West, Zimbabwe: a cluster-randomised controlled trial

Gabida, M.; Kazonga, E.; Bowa, K.

2026-08-31 public and global health 10.64898/2026.08.26.26361392 medRxiv
Top 2%
0.0%
Show abstract

Abstract Preventable neonatal deaths remain a major public health problem in Zimbabwe, where near-universal antenatal and facility-delivery coverage coexist with a rising neonatal mortality rate. This study evaluated whether institutionalising three core "vital signs" of the community health system (a trained village health worker (VHW) workforce, functional community governance structures, and modified women's and men's participatory learning and action groups) reduces preventable neonatal deaths in Mashonaland West Province. An embedded QUAN (qual) mixed-methods design was used, with a two-arm, parallel-group cluster-randomised controlled trial as the dominant strand. Fifty-two ward-level clusters were randomised 1:1 to the institutionalised community health system package or to standard Ministry of Health and Child Care community services, and 984 pregnant women were enrolled between 1 September 2020 and 31 October 2021, with each mother-infant pair followed to 28 days after delivery, yielding 973 mother-infant pairs for intention-to-treat analysis. The primary outcome was neonatal death within 28 days of life, expressed per 1,000 live births. The primary analysis used a three-level mixed-effects log-binomial regression model with cluster and community-health-worker random intercepts, adjusted for pre-specified covariates. Supervised machine-learning classifiers with leave-one-cluster-out cross-validation, Cox proportional-hazards regression, and multilevel logistic models were fitted as supplementary analyses. An embedded longitudinal process evaluation used key informant interviews and focus group discussions, which were analysed thematically and integrated with the quantitative findings. The neonatal mortality rate was 44.8 per 1,000 live births in the intervention arm versus 110.1 per 1,000 in the control arm. The adjusted risk ratio for neonatal death was 0.43 (95% CI 0.26-0.70; p < 0.001), a 57% relative reduction, with a number needed to treat of 16 mother-infant pairs (95% CI 11-29). Low birthweight (<2,500 g), birth interval under two years, and low community women's literacy were the strongest risk factors, while trained VHWs, functional community governance, early antenatal care, and sustained participatory group attendance were independently protective. The women's and men's groups were protective in a dose-dependent manner, becoming significant at four or more cycles (about 14 meetings) (adjusted odds ratio 0.71; 95% CI 0.60-0.85; p = 0.001). A random forest classifier discriminated against neonatal deaths with a cross-validated area under the curve of 0.842 and a sensitivity of 0.912. Qualitative findings converged with the trial results, identifying male engagement, earlier care-seeking, danger-sign literacy, social-network activation, and community death audits as the behavioural and structural mechanisms of change. Institutionalising the community health system package (trained VHWs, functional governance, early antenatal engagement, and sustained participatory groups) was associated with a substantial reduction in preventable neonatal deaths. The findings suggest that in high-coverage, high-mortality settings, the binding constraint is structural rather than clinical, and that scaling functional community governance and workforce infrastructure in the most disadvantaged communities may accelerate progress toward neonatal survival targets. The principal limitations are a one-year follow-up period, the rarity of neonatal death, and concurrent national programming that only partially reached the control clusters. Trial registration: Pan African Clinical Trials Registry, PACTR202607591142118 (https://pactr.samrc.ac.za/TrialDisplay.aspx?TrialID=PACTR202607591142118); registered retrospectively on 7 July 2026.

12
Multi-organ aging quantified from routine chest CT predicts chronic disease risk and mortality

Sato, J.; Salehjahromi, M.; Zafar, A.; Muneer, A.; Xu, X.; Zhu, E.; Vokes, N. I.; Cascone, T.; Le, X.; Altan, M.; Gardner, E. E.; Sheshadri, A.; Ostrin, E. J.; Salahudeen, A. A.; Li, T.; Merad, M.; Chaudhuri, A. A.; Gerber, D. E.; Kay, F. U.; Godoy, M. C. B.; Carter, B. W.; Shroff, G. S.; Byers, L. A.; Chung, C.; Jaffray, D.; Rice, D.; Liao, Z.; Chang, J. Y.; Vaporciyan, A. A.; Gibbons, D. L.; Wu, C. C.; Heymach, J. V.; Zhang, J.; Wu, J.

2026-08-31 radiology and imaging 10.64898/2026.08.26.26361434 medRxiv
Top 2%
0.0%
Show abstract

Biological aging occurs heterogeneously across individuals and organs. However, current measures of biological age incompletely capture organ-specific differences in health and disease risk. Because chest CT visualizes multiple thoracic organs, it offers an opportunity to quantify structural aging across organ systems. Here, we developed MOSAIC-Age, a framework characterizing eight organ-specific aging clocks on chest CT. The clocks were developed and validated using 9,971 CT scans from CT-RATE and MIDRC, and subsequently locked and applied to two independent prospective cohorts with 35,293 participants from the National Lung Screening Trial and Genetic Epidemiology of COPD study. CT-derived biological age gaps (BAGs) were examined in relation to lifestyle and socioeconomic factors, prevalent comorbidities, incident chronic diseases, and all-cause and cause-specific mortality. Higher BAGs, indicating organs that appeared older on CT than expected for their chronological age, were broadly associated with adverse health characteristics, chronic disease burden, and increased mortality risk. Multiple disease outcomes were associated with aging across several organs, whereas in multivariable analyses including all eight organ-specific BAGs, the remaining associations were more organ specific. A greater number of markedly older-appearing organs and a faster pace of aging were each associated with higher mortality. Together, these findings demonstrate that routine chest CT captures both shared and organ-specific patterns of biological aging and establish CT-derived organ aging as a quantitative imaging biomarker for assessing multi-organ health and long-term disease risk.

13
Clinically Generalisable End-to-End Graph Learning for CT Image-Based Multitask Stroke Diagnosis

Lu, Z.; Uddin, S.; Uribe, S.; White, S.; Martins, R. T.; Chau, S.; Mosaddek, A. S. M.; Islam, M. S.; Nahar, N.; Azad, A. K. M.; Hossain, K. M. N.; Choudhury, H. S.; Hasan, K. M. R.; Mosaddek, N.; Rahman, S.; Hossain, M. M.; Sizar, K. M. M. H.; Angione, C.; Lio, P.; Islam, M. T.; Moni, M. A.

2026-08-31 radiology and imaging 10.64898/2026.08.26.26360026 medRxiv
Top 2%
0.0%
Show abstract

Stroke remains a leading cause of mortality and long-term disability worldwide, yet rapid diagnosis is often limited by the shortage of trained radiologists, particularly in resource-constrained settings. Automated analysis of CT imaging offers a potential solution, but existing methods often struggle to achieve clinically generalisable performance while jointly addressing multiple diagnostic tasks. Here we present the Intelligent Integrated Stroke Diagnosis System IISDS, an end-to-end deep learning framework built upon StrokeGNN, a graph-based architecture that integrates 3D contextual feature extraction with U-Net-based 2D lesion segmentation to enable comprehensive stroke analysis from non-contrast CT scans. IISDS performs stroke subtype classification, lesion segmentation and lesion volume estimation within a unified pipeline. To develop and validate the system, we collected and curated BGD-ISD through a collaboration between AI researchers, neurologists, radiologists and clinicians, resulting in a large multi-centre dataset comprising 1,507 CT scans from 597 stroke cases acquired across six hospitals and medical centres in Bangladesh. Across BGD-ISD and multiple publicly available datasets, IISDS achieves state-of-the-art performance on all tasks, improving segmentation accuracy by [&ge;]0.011 Dice score, reducing lesion volume estimation error by [&ge;]0.3 average symmetric surface distance (ASSD), and increasing classification performance by [&ge;]0.018 area under the receiver operating characteristic curve (AUC) compared with existing approaches. These results demonstrate the potential of graph-based deep learning to enable clinically generalisable, automated and scalable stroke diagnosis from CT imaging, supporting rapid clinical decision-making, particularly in healthcare environments with limited access to expert radiological interpretation.

14
Patient factors influencing participation in aerobic exercise during inpatient and outpatient rehabilitation post stroke: a prospective cohort study

Barzideh, A.; Devasahayam, A. J.; Marzolini, S.; Munce, S.; Sibley, K. M.; Inness, E. L.; Mansfield, A.

2026-08-31 rehabilitation medicine and physical therapy 10.64898/2026.08.26.26361451 medRxiv
Top 2%
0.0%
Show abstract

Background: Aerobic exercise is recommended during stroke rehabilitation to improve cardiorespiratory fitness and support recovery; however, participation rates remain low. While institutional and system-level barriers have been widely examined, less is known about how individual patient factors influence engagement in aerobic exercise during rehabilitation. Objectives: We aimed to determine whether depressive symptoms, apathy, self-efficacy and outcome expectations for exercise, perceived barriers, or past exercise history were associated with aerobic exercise participation in stroke rehabilitation. Methods: In this prospective cohort sub-study, adults admitted to in- or out-patient stroke rehabilitation at three urban hospitals completed validated questionnaires assessing depressive symptoms, apathy, exercise self-efficacy, outcome expectations for exercise, perceived barriers to being active, and premorbid exercise history. Participants were separated into two groups for analysis: those who completed aerobic exercise during rehabilitation and those who did not. Equivalence testing and between-group comparisons were performed. Results: Sixty-two participants were enrolled; 16 participated in aerobic exercise and 46 did not. Groups were not equivalent on any individual-level factors. Compared to non-participants, those who performed aerobic exercise had significantly higher depressive symptom scores (p=0.0025) and lower self-efficacy for exercise (p=0.0087). Non-participants demonstrated significantly higher apathy (p=0.0007). No significant differences were found for outcome expectations, perceived barriers, or exercise history. Conclusion: Depressive symptoms and lower self-efficacy did not impede aerobic exercise participation during rehabilitation. Increased apathy, however, was associated with non-participation. Findings highlight the need for individually tailored aerobic exercise prescriptions that consider motivational and affective factors to optimize engagement during stroke rehabilitation.

15
Burr-Hole Intersection of Middle Meningeal Artery Branches and Recurrence in Chronic Subdural Haematoma: a Multicentre Retrospective Cohort Study

Saba, T. M.; Moudgil-Joshi, J.; Pandit, A. S.; Penn, J.; Mallon, D.; Marcus, H. J.; Grover, P.

2026-08-31 surgery 10.64898/2026.08.26.26361348 medRxiv
Top 2%
0.0%
Show abstract

Background and Objectives: Recurrence following burr-hole drainage of chronic subdural haematoma (cSDH) occurs in 10-25% of cases, sustained by neovascularisation of the subdural neomembrane supplied by the middle meningeal artery (MMA). MMA embolisation reduces recurrence; whether incidental burr-hole intersection of MMA branches during drainage confers similar benefit is unknown. Methods: We performed a multicentre retrospective cohort study of consecutive adults undergoing burr-hole drainage for cSDH at two UK tertiary neurosurgical centres. Postoperative thin-slice CT was used to classify burr-hole intersection of the underlying MMA groove (no hit, distal-branch hit or main-branch hit) and measure perpendicular burr-hole-to-MMA-groove distance. Co-primary outcomes were radiological recurrence and recurrence requiring intervention. Patient-clustered multivariable logistic regression adjusted for prespecified clinical covariates and treating site. Results: 227 patients (284 operated hemispheres) were included. Radiological recurrence decreased from 34.4% with no branch hit to 22.9% with main-branch intersection, with the gradient confined predominantly to unilateral cSDH. Main-branch intersection was associated with lower adjusted odds of radiological recurrence in unilateral cSDH (adjusted OR 0.30, 95% CI 0.11- 0.81; P = .018), with a similar but non-significant association in the overall cohort (adjusted OR 0.53, 95% CI 0.26-1.07; P = .075). Burr-hole-to-MMA-groove distance demonstrated a more consistent association: in the overall cohort, each 5-mm increase independently increased the odds of radiological recurrence (adjusted OR 1.38, 95% CI 1.04-1.82; P = .025). In unilateral cSDH, each 5-mm increase was independently associated with both radiological recurrence (adjusted OR 1.45, 95% CI 1.03-2.04; P = .034) and recurrence requiring intervention (adjusted OR 1.52, 95% CI 1.05-2.20; P = .027). Conclusion: Main-branch intersection of the middle meningeal artery during routine burr-hole surgery is associated with lower recurrence of unilateral cSDH, while the accompanying burr-hole-to-MMA-groove distance gradient provides biologically plausible support for a dose-response relationship. Together, these findings provide mechanistic rationale for prospective evaluation of intentional neuronavigation-guided MMA targeting (BURR-MMA; NCT07549893).

16
Preferences for receiving study results among pregnant women participating in a phase III clinical trial in Papua New Guinea.

Mengi, A.; Bagita-Vangana, M.; Tesine, P.; Laman, M.; Bolnga, J. W.; Ome-Kaius, M.; Kulimbao, J.; Mase, J.; Mal, L. S.; Mnjala, H.; Lee, G.; Cassidy-Seyoum, S. A.; Thriemer, K.; Unger, H. W.

2026-08-31 medical ethics 10.64898/2026.08.27.26361571 medRxiv
Top 2%
0.0%
Show abstract

Disseminating study results to participants is an ethical responsibility for researchers but remains uncommon in low- and middle-income countries, and participants preferences for receiving study results are poorly understood. This study examined study result dissemination preferences among pregnant women in a phase III malaria prevention trial in Papua New Guinea (PNG). Participants completed an interviewer-administered questionnaire (survey) assessing their interest in and motivation for receiving trial results and preferences for dissemination methods and content. Associations between participants characteristics and dissemination preferences were explored using multivariable logistic regression analysis. Of 1172 trial participants, 96.0% (1125/1172) completed the survey, and of these 99.6% (1121/1125) wanted to learn about the trial results. The main motivation factors driving participants interest were an acknowledgment of their contribution to research (51.7%; n=579) and a better understanding of the study (45.0%; n=505). Most participants (78.9%; n=884) wanted to learn about the trial findings through written summary and a group meeting with other participants at the nearest clinic (31.1%, n=349). Multivariable regression analysis indicated that participants from rural/peri-urban clinics were more likely to choose non-electronic media dissemination approaches such as a group meeting as compared to urban-dwelling participants. Frequently selected items (>50% of participants) for content included information regarding good results of the study, purpose of the study, medical treatment advances, results specific to me, and how study was conducted. There was heterogenicity in the preference for dissemination content: compared to urban clinics rural clinics are less likely to want to learn about how and why study was conducted and medical and scientific advances. Overall, the majority wanted to learn about trial results, highlighting the importance of integrating dissemination into research activities in PNG. Variation in preferences for mode and content of dissemination between study clinics suggests that dissemination activities could be tailored to local context and preferences.

17
Paradoxical relief after seizures: a diagnostic signal distinguishing functional/dissociative from epileptic seizures

Masharani, A.; Koreki, A.; Marcelo, M.; Shalfrooshan, K.; Diamos, M.-A.; Santucci, C.; Pillai, K.; Bindman, D.; O'Sullivan, S.; Rugg-Gunn, F.; Sidhu, M.; Yogarajah, M.

2026-08-31 neurology 10.64898/2026.08.27.26360607 medRxiv
Top 2%
0.0%
Show abstract

Objective: To determine whether paradoxical relief, feeling unusually better after a seizure compared to before it, is more common after functional/dissociative seizures (FDS) than epileptic seizures (ES), quantify its diagnostic accuracy, and explore its relationship with preictal symptoms. Methods: Consecutive patients admitted to a tertiary epilepsy unit for prolonged inpatient EEG monitoring underwent a structured clinical interview on admission, before final multidisciplinary diagnostic classification. Preictal dissociative and autonomic/somatic symptom burden was assessed using items adapted from established questionnaires. Diagnostic classification incorporated clinical history, seizure semiology, video electroencephalography findings, and collateral information. Patients with dual or indeterminate diagnoses were excluded. Associations with paradoxical relief were examined using logistic regression, followed by an exploratory mediation analysis. Results: Of 176 patients assessed, 66 with FDS and 65 with ES were included. Paradoxical relief was reported by 46/66 patients with FDS (69.7%) and 10/65 with ES (15.4%; unadjusted odds ratio [OR] 12.65, 95% confidence interval [CI] 5.57 to 31.09). As a diagnostic signal for FDS, paradoxical relief had 69.7% sensitivity (95% CI 57.1 to 80.4), 84.6% specificity (95% CI 73.5 to 92.4), a positive likelihood ratio of 4.53 (2.51 to 8.19), and a negative likelihood ratio of 0.36 (0.24 to 0.52). FDS diagnosis remained independently associated with paradoxical relief after adjustment (OR 10.59, 95% CI 3.42 to 38.06). In a parallel mediation analysis, dissociative symptom burden showed a significant indirect effect, accounting for 19.5% of the association between diagnostic group and relief, whereas the indirect effect through somatic/autonomic symptom burden was not significant. Significance: Paradoxical relief is substantially more common after FDS than ES and may provide a simple, clinically useful diagnostic signal. Its absence does not exclude FDS, and the finding requires external validation. The association with dissociative symptoms is exploratory and supports prospective investigation of whether relief reflects transient resolution of a disturbed, disembodied preictal state.

18
Urinary collagen type I degradation products as common fibrosis biomarkers in chronic diseases

Mina, I. K.; Hussain, Y.; Siwy, J.; Catanese, L.; Rupprecht, H.; Beige, J.; Staessen, J. A.; Metzger, J.; Persson, F.; Rossing, P.; Delles, C.; Schanstra, J. P.; Bannaga, A.; Vlahou, A.; Mischak, H.; Arasaradnam, R. P.; Latosinska, A.

2026-08-31 nephrology 10.64898/2026.08.26.26361420 medRxiv
Top 2%
0.0%
Show abstract

Background: Fibrosis, characterised by excessive accumulation of collagen type I (COL1), is a common feature of chronic diseases, including liver diseases (LDs), chronic kidney disease (CKD) and heart failure (HF). COL1 degradation products can be detected in urine by proteomics/ peptidomics analyses and may serve as non-invasive biomarkers of fibrosis. We aimed to identify a common molecular signature of fibrosis across these diseases that may ultimately guide interventions to slow disease progression and prevent organ damage. Methods: Using capillary electrophoresis coupled to mass spectrometry (CE-MS), naturally occurring COL1 degradation products (peptides) in the urine of patients with fibrotic disease, LDs (n=127), CKD (n=263) or HF (n=187), were investigated and compared with the same number of matched controls. Disease-associated COL1 peptides were identified separately for each condition, and peptides showing consistent associations across the three diseases were selected to define a common fibrosis signature. A support vector machine model based on the selected peptides was developed and validated in independent cohorts of patients with LDs (n=110), CKD (n=93), HF (n=32) and controls (n=643). Results: We identified a common fibrotic signature consisting of 50 COL1 degradation products, mainly downregulated in fibrosis. A model based on these peptides achieved a strong performance, with an area under the receiver operating characteristic curve (AUC) of 0.935 (95% confidence interval (CI) 0.917-0.953, p<0.0001) in an external validation cohort comprising pooled disease groups (LDs, CKD, and HF) and controls. Performance was maintained in LDs, CKD and HF, with AUCs of 0.917 (95% CI 0.890-0.944, p<0.0001), 0.951 (95% CI 0.931-0.971, p<0.0001) and 0.950 (95% CI 0.903-0.997, p<0.0001), respectively. The model scores were significantly associated with fibrosis stage in LDs (p=0.0097) and with interstitial fibrosis and tubular atrophy in CKD (p=0.045). Conclusion: A model of urinary COL1 peptides captures a shared collagen degradation signature across organs and diseases, enabling the non-invasive assessment of fibrosis irrespective of its origin. As these peptides exclusively reflect collagen degradation, the findings suggest impaired collagen degradation as a driver in fibrosis. Future clinical studies are warranted to evaluate the utility of this model for early fibrosis detection and earlier implementation of anti-fibrotic interventions.

19
Acceptability, feasibility, quality of life and diabetes distress score outcomes: A pragmatic randomised clinical trial on continuous glucose monitoring for people with type 1 diabetes

Marban-Castro, E.; Muhwava, L.; Girdwood, S.; Kemp, T.; Freitas, J.; Kamau, Y.; Otieno, M.; Akach, D.; Morato, A.; Sanz, S.; Fiechter, V.; Erkosar, B.; Watson, M.; Vetter, B.; Haldane, C.; Shilton, S.; Rheeder, P.; Dave, J. A.; Carrihill, M.; Karsas, M.

2026-08-31 endocrinology 10.64898/2026.08.26.26361479 medRxiv
Top 2%
0.0%
Show abstract

Introduction: Continuous glucose monitoring (CGM) offers an advancement over traditional self-monitoring of blood glucose (SMBG) for people living with type 1 diabetes (T1D). However, evidence on the acceptability and feasibility of different CGM use cases in African populations remains limited. Methods: This was a pragmatic three-arm, randomised controlled trial on CGM conducted among people living with T1D in three public healthcare clinics in South Africa. Participants were assigned to Arm 1 (continuous CGM), Arm 2 (periodic CGM), or Arm 3 (SMBG). Diabetes education was provided at all study visits. Feasibility was assessed by adherence to CGM use and through the Glucose Monitoring Satisfaction Survey (GMSS). Diabetes distress was measured by the Diabetes Distress Scale (DDS), health-related quality of life (HRQoL) by the EQ-5D scales, and acceptability using the Theoretical Framework of Acceptability (TFA). Surveys were collected on paper and transferred to OpenClinica. Analyses were performed in R. The trial was registered in the Clinical Trials Registry (NCT05944718) on July 13, 2023. Results: A total of 83 participants were included in Arm 1, 85 in Arm 2, and 80 in Arm 3. CGM mean active time was 55% in Arm 1 versus 69% in Arm 2. The proportion of participants meeting the [&ge;]70% active time threshold was higher in Arm 2 (52%) than in Arm 1 (34%). Diabetes' distress declined across arms during the intervention period, with no significant difference between arms; distress increased slightly six months post-intervention but remained below baseline. At 6 months, glucose monitoring satisfaction was significantly higher in both CGM arms than in the SMBG arm, and satisfaction increased over time in CGM arms. Health-related quality of life remained stable across arms during the intervention period with no significant difference between arms. High acceptability was observed in both CGM arms, with higher ratings in the periodic arm. Conclusions: CGM was acceptable to people living with type 1 diabetes and feasible to use in public-sector clinics in South Africa, with high acceptability under continuous and periodic use. Health-related quality of life remained stable across arms, and diabetes-related distress declined, during the intervention period, across arms. Glucose monitoring satisfaction rose significantly in both CGM arms compared to SMBG. Periodic CGM might be a promising and potentially more scalable option than continuous use for public-sector care.

20
Antiseizure Medication Administration Gaps Across the ICU-to-Floor Transfer: A Matched Within-Patient Comparison

Gorenshtein, A.; Adiniaev, Y.; Srour, A.; Klang, E.; Daniel, O.

2026-08-31 neurology 10.64898/2026.08.26.26361462 medRxiv
Top 2%
0.0%
Show abstract

Objective: Whether a scheduled antiseizure medication (ASM) continues on schedule across the ICU-to-floor transfer has not been characterized. We quantified ASM administration-gap frequency across this transfer and compared it with gap frequency during matched non-transfer intervals in the same patient and drug. Methods: In this retrospective MIMIC-IV (version 3.1) cohort study, we identified epilepsy and status-epilepticus admissions with an ICU stay followed by floor transfer and a scheduled ASM order active at ICU departure. A gap was defined as an interval exceeding 1.5 times the expected dosing interval between the last ICU dose and first floor dose, or no further dose before discharge, and compared with a matched non-transfer control interval in the same patient and drug (paired McNemar test). A multivariable model evaluated six prespecified clinical predictors; sociodemographic variables were summarized descriptively. Results: Among 2,469 ASM transition-by-drug observations (1,583 admissions, 1,335 patients), an administration gap occurred in 251 (10.2%; 95% CI, 8.7%-11.7%). Gap frequency across the transfer exceeded frequency during matched non-transfer control intervals in the same patient and drug: a paired rate difference of 5.8 percentage points (95% CI, 4.4-7.1; 7.5% vs 1.7%; P = 7.3 x 10^-22) before the transfer and 6.4 percentage points (95% CI, 4.9-7.9; 8.9% vs 2.5%; P = 1.9 x 10^-23) after. Gap rates were similar for intravenous-available (9.9%) and oral-only (11.4%) drugs (rate difference, 1.5 percentage points; 95% CI, -1.6 to 4.5; P = .34). None of six prespecified predictors reached significance after correction. Significance: An antiseizure medication administration gap occurred in approximately 1 of every 10 drug-transition observations at the ICU-to-floor transfer, exceeding matched non-transfer gap rates by 5.8 to 6.4 percentage points. This transfer-associated excess, rather than any single medication or patient characteristic, supports a structured medication-continuity check.